Papers with human well-being

    1 papers
    Improving Generalizability in Implicitly Abusive Language Detection with Concept Activation Vectors (2022.acl-long)

    Copied to clipboard

    Challenge: a new study shows that general abusive language classifiers are reliable in detecting explicit abuse but fail to detect more subtle abuses.
    Approach: They propose an interpretability technique to quantify the sensitivity of a trained model to new data . they propose a degree of explicitness metric to suggest out-of-domain unlabeled examples .
    Outcome: The proposed interpretability technique is useful for predicting the generalizability of the model on new data.

    What is GenGO?

    GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

    Information

    About
    Limitations